Papers with annotation layer
Placing multi-modal, and multi-lingual Data in the Humanities Domain on the Map: the Mythotopia Geo-tagged Corpus (2022.lrec-1)
Copied to clipboard
Voula Giouli, Anna Vacalopoulou, Nikolaos Sidiropoulos, Christina Flouda, Athanasios Doupas, Giorgos Giannopoulos, Nikos Bikakis, Vassilis Kaffes, Gregory Stainhaouer
| Challenge: | Using mythology as a starting point, visitors of Northern Greece will have a multi-faceted experience using a corpus of textual data supplemented with images, and video. |
| Approach: | They propose to integrate a multi-lingual corpus with a dedicated database with advanced indexing, linking and search functionalities into a platform for scholarly research in the digital humanities. |
| Outcome: | The proposed infrastructure will be integrated into a platform aimed at providing a multi-faceted experience to visitors of Northern Greece using mythology as a starting point. |
Cross-type French Multiword Expression Identification with Pre-trained Masked Language Models (2024.lrec-main)
Copied to clipboard
| Challenge: | Multiword expressions (MWEs) have linguistic features that distinguish them from regular word groupings. |
| Approach: | They propose a combination of two systems that learn verbal multiword expressions and non-verbal MWEs to improve performance on a cross-type dataset . |
| Outcome: | The proposed system improves the F1 score on a french treebank with VMWEs and nVMWES training data. |